Papers with converted encoder models
Look Both Ways and No Sink: Converting LLMs into Text Encoders without Training (2025.acl-long)
Copied to clipboard
| Challenge: | Existing methods for converting large language models into powerful text encoders require extensive training on large datasets. |
| Approach: | They propose a training-free approach that enables bidirectional attention and suppresses the attention sink phenomenon, resulting in superior performance. |
| Outcome: | The proposed approach enables bidirectional attention and suppresses the attention sink phenomenon, resulting in superior performance. |